Temporal Synchronization and Normalization of Speech Videos for Face Recognition

نویسندگان

Usman Saeed

Jean-Luc Dugelay

چکیده

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Lip Analysis for Person Recognition

The human face is an attractive biometric identifier and face recognition has certainly improved a lot since its beginnings some three decades ago, but still its application in real world has achieved limited success. In this doctoral dissertation we focus on a local feature of the human face namely the lip and analyse it for its relevance and influence on person recognition. In depth study is ...

متن کامل

Speech Emotion Recognition Using Scalogram Based Deep Structure

Speech Emotion Recognition (SER) is an important part of speech-based Human-Computer Interface (HCI) applications. Previous SER methods rely on the extraction of features and training an appropriate classifier. However, most of those features can be affected by emotionally irrelevant factors such as gender, speaking styles and environment. Here, an SER method has been proposed based on a concat...

متن کامل

Improving the performance of MFCC for Persian robust speech recognition

The Mel Frequency cepstral coefficients are the most widely used feature in speech recognition but they are very sensitive to noise. In this paper to achieve a satisfactorily performance in Automatic Speech Recognition (ASR) applications we introduce a noise robust new set of MFCC vector estimated through following steps. First, spectral mean normalization is a pre-processing which applies to t...

متن کامل

Effects of ageing on speed and temporal resolution of speech stimuli in older adults

Background: According to previous studies, most of the speech recognition disorders in older adults are the results of deficits in audibility and auditory temporal resolution. In this paper, the effect of ageing on timecompressed speech and auditory temporal resolution by word recognition in continuous and interrupted noise was studied. Methods: A time-compressed speech test (TCST) w...

متن کامل

Visual speech influences speeded auditory identification

Auditory speech perception is faster and more accurate when combined with visual speech. We attempted to replicate previous findings that suggested visual speech facilitates auditory processing when speech is paired with matching video and interferes with processing when paired with mismatched videos. Crucially we employed button presses instead of a vocal response to determine if previous resu...

متن کامل

ذخیره در منابع من

ذخیره در منابع من قبلا به منابع من ذحیره شده

{@ msg_add @}

با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره شماره

صفحات -

تاریخ انتشار 2017

Temporal Synchronization and Normalization of Speech Videos for Face Recognition

نویسندگان

چکیده

منابع مشابه

Lip Analysis for Person Recognition

Speech Emotion Recognition Using Scalogram Based Deep Structure

Improving the performance of MFCC for Persian robust speech recognition

Effects of ageing on speed and temporal resolution of speech stimuli in older adults

Visual speech influences speeded auditory identification

عنوان ژورنال:

اشتراک گذاری